Papers with India Radio

1 papers
Textless Speech-to-Speech Translation With Limited Parallel Data (2024.findings-emnlp)

Copied to clipboard

Challenge: Existing speech-to-speech translation models either leverage text as an intermediate step or require hundreds of hours of parallel speech data.
Approach: They propose a framework for training textless S2ST models that require dozens of hours of parallel speech data.
Outcome: The proposed model achieves reasonable performance on three domains with single-speaker synthesized speech.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations